Search results

1 – 9 of 9
Article
Publication date: 10 April 2023

Evagelos Varthis and Marios Poulos

This study aims to present metaGraphos, a crowdsourcing system that aids in the transcription and semantic enhancement of scanned documents by using a pool of volunteers or people…

Abstract

Purpose

This study aims to present metaGraphos, a crowdsourcing system that aids in the transcription and semantic enhancement of scanned documents by using a pool of volunteers or people willing to participate in exchange for a financial reward.

Design/methodology/approach

The metaGraphos can be used in circumstances where optical character recognition fails to produce satisfactory results, semantic tagging or assigning thematic headings to texts is considered necessary or even when ground-truth data has to be collected in raw form.

Findings

The system automatically provides a Web-based interface comprising a static HTML page and JavaScript code that displays the scanned images of the document, coupled with the corresponding incomplete texts side by side, allowing users to correct or complete the texts in parallel.

Social implications

By assisting the parallel transcription and the semantic enhancement of difficult scanned documents, the system further reveals the hidden cultural wealth and aids in knowledge dissemination, a fact that contributes significantly to the academic-scientific dialog and feedback.

Originality/value

Individual researchers, libraries and organizations in general may benefit from the system because it is cost-effective, practical and simple to set up client–server architecture that provides a reliable way to transcribe texts or revise transcriptions on a large scale.

Details

Collection and Curation, vol. 42 no. 4
Type: Research Article
ISSN: 2514-9326

Keywords

Article
Publication date: 19 May 2021

Evagelos Varthis, Marios Poulos, Ilias Giarenis and Sozon Papavlasopoulos

This study aims to provide a system capable of static searching on a large number of unstructured texts directly on the Web domain while keeping costs to a minimum. The proposed…

Abstract

Purpose

This study aims to provide a system capable of static searching on a large number of unstructured texts directly on the Web domain while keeping costs to a minimum. The proposed framework is applied to the unstructured texts of Migne’s Patrologia Graeca (PG) collection, setting PG as an implementation example of the method.

Design/methodology/approach

The unstructured texts of PG have automatically transformed to a read-only not only Structured Query Language (NoSQL) database with a structure identical to that of a representational state transfer access point interface. The transformation makes it possible to execute queries and retrieve ranked results based on a specialized application of the extended Boolean model.

Findings

Using a specifically built Web-browser-based search tool, the user can quickly locate ranked relevant fragments of texts with the ability to navigate back and forth. The user can search using the initial part of words and by ignoring the diacritics of the Greek language. The performance of the search system is comparatively examined when different versions of hypertext transfer protocol (Http) are used for various network latencies and different modes of network connections. Queries using Http-2 have by far the best performance, compared to any of Http-1.1 modes.

Originality/value

The system is not limited to the case study of PG and has a generic application in the field of humanities. The expandability of the system in terms of semantic enrichment is feasible by taking into account synonyms and topics if they are available. The system’s main advantage is that it is totally static which implies important features such as simplicity, efficiency, fast response, portability, security and scalability.

Details

International Journal of Web Information Systems, vol. 17 no. 3
Type: Research Article
ISSN: 1744-0084

Keywords

Article
Publication date: 16 August 2021

Evagelos Varthis, Spyros Tzanavaris, Ilias Giarenis, Sozon Papavlasopoulos, Manolis Drakakis and Marios Poulos

This paper aims to present a methodology for the semantic enrichment on the scanned collection of Migne’s Patrologia Graeca (PG), attempting to easily locate on the Web domain the…

Abstract

Purpose

This paper aims to present a methodology for the semantic enrichment on the scanned collection of Migne’s Patrologia Graeca (PG), attempting to easily locate on the Web domain the scanned PG source, when a reference of this source is described and commented on another scanned or textual document, and to semantically enrich PG through related scanned or textual documents named “satellite texts” published by third people. The present enrichment of PG uses as satellite texts the Dorotheos Scholarios's Synoptic Index (DSSI) which act as metadata for PG.

Design/methodology/approach

The methodology consists of two parts. The first part addresses the DSSI transcription via a proper web tool. The second part is divided into two subsections: the accomplishment of interlinking the printed column numbers of each scanned PG page with its actual filename, which is the build of a matching function, and the build of a web interface for PG, based on the generated Uniform Resource Identifiers (URIs) of the above first subsection.

Findings

The result of the implemented methodology is a Web portal, capable of providing server-less search of topics with direct (single click) navigation to sources. The produced system is static, scalable, easy to be managed and requires minimal cost to be completed and maintained. The produced data sets of transcribed DSSI and the JavaScript Object Notation (JSON) matching functions are available for personal use of students and scholars under Creative Commons license (CC-BY-NC-SA).

Social implications

Scholars or anyone interested in a particular subject can easily locate topics in PG and reference them, using URIs that are easy to remember. This fact contributes significantly to the related scientific dialogue.

Originality/value

The methodology uses the transcribed satellite texts of DSSI, which act as metadata for PG, to semantically enrich PG collection. Furthermore, the built PG Web interface can be used by other satellite texts as a reference basis to further enrich PG, as it provides a direct identification of sources. The presented methodology is general and can be applied to any scanned collection using its own satellite texts.

Details

Information Discovery and Delivery, vol. 50 no. 2
Type: Research Article
ISSN: 2398-6247

Keywords

Article
Publication date: 27 September 2011

Marios Poulos, Nikolaos Korfiatis and George Bokos

This paper aims to present the semantic content identifier (SCI), a permanent identifier, computed through a linear‐time onion‐peeling algorithm that enables the extraction of…

Abstract

Purpose

This paper aims to present the semantic content identifier (SCI), a permanent identifier, computed through a linear‐time onion‐peeling algorithm that enables the extraction of semantic features from a text, and the integration of this information within the permanent identifier.

Design/methodology/approach

The authors employ SCI to propose a mechanism for simultaneously checking the authenticity and degrees of similarity between different information objects, and present an empirical investigation of the method. A management scenario for the control of the authentication process and the detection of the degree of violation of documents is proposed.

Findings

Such a mechanism could be adopted as a component of libraries' strategy for the protection of the copyrights for documents published on the web.

Practical implications

The use of the proposed numeric code can be utilised efficiently as a constituent part of the digital object identifier (DOI) system, making its computation more efficient and meaningful.

Originality/value

The identifier proposed in the paper can result in a more efficient index for identifying and retrieving objects in a digital library, as well as online repositories and commercial applications that can handle information retrieval requests more effectively.

Article
Publication date: 17 April 2007

Nikolaos Korfiatis, Marios Poulos and George Bokos

The purpose of this research is to address the need for a definition of metadata descriptors for use in enhancing the accuracy of bibliometric instruments of scholarly evaluation…

1154

Abstract

Purpose

The purpose of this research is to address the need for a definition of metadata descriptors for use in enhancing the accuracy of bibliometric instruments of scholarly evaluation, such as the impact factor.

Design/methodology/approach

A semantic vocabulary – COAP – is constructed, deployed on top of the Resource Description Framework (RDF), by extending the Friend‐of‐a‐Friend (FOAF) schema.

Findings

An extension of the FOAF vocabulary is considered as the ability to describe a publication record such as this paper in terms of scholar contributions and participations. In order to achieve that, the FOAF vocabulary is extended.

Practical implications

The application of this semantic vocabulary could be used as a way of enhancing the accuracy of source data for bibliometric evaluation instruments.

Originality/value

The paper discusses how metadata descriptors can contribute to the improvement of already established scholar evaluation instruments such as the impact factor. It will be of use in the development of digital libraries.

Details

The Electronic Library, vol. 25 no. 2
Type: Research Article
ISSN: 0264-0473

Keywords

Article
Publication date: 11 May 2012

Sozon Papavlasopoulos and Marios Poulos

Over the past decades scientific advances have inspired major technological innovations in academic libraries. Academic libraries have transformed into information centers of new…

Abstract

Purpose

Over the past decades scientific advances have inspired major technological innovations in academic libraries. Academic libraries have transformed into information centers of new quality, becoming thus an integral part of an academic institution's teaching and research curriculum. Assessing a library's functions and services has become an imperative need. The scope of this study, therefore, is to define a theoretical model for the combination of all individual assessment indicators into one single number‐indicator.

Design/methodology/approach

The Monte Carlo technique is suggested in order to deal with the problems of the objectivity expert's decision and the collection of a large number of indicators. The methodology used is articulated into four stages: the normalization of single values; the setting of the indicator weighting according to the expert's opinion; the construction of a well‐fitted neural network; and the training and testing procedure of the neural network.

Findings

The proposed method solves problems related to the normalization of measured data linked with significant relevant properties of the library services evaluation.

Originality/value

The practice of using a neural network in evaluating library services as introduced by this paper is characterized as novel.

Article
Publication date: 1 May 2006

Nikolaos Th. Korfiatis, Marios Poulos and George Bokos

The purpose of this paper is to present an approach to evaluating contributions in collaborative authoring environments, and in particular, Wikis using social network measures.

7918

Abstract

Purpose

The purpose of this paper is to present an approach to evaluating contributions in collaborative authoring environments, and in particular, Wikis using social network measures.

Design/methodology/approach

A social network model for Wikipedia has been constructed, and metrics of importance such as centrality have been defined. Data has been gathered from articles belonging to the same topic using a web crawler, in order to evaluate the outcome of the social network measures in the articles.

Findings

Finds that the question of the reliability regarding Wikipedia content is a challenging one and as Wikipedia grows, the problem becomes more demanding, especially for topics with controversial views such as politics or history.

Practical implications

It is believed that the approach presented here could be used to improve the authoritativeness of content found in Wikipedia and similar sources.

Originality/value

This work tries to develop a network approach to the evaluation of Wiki contributions, and approaches the problem of quality Wikipedia content from a social network point of view.

Details

Online Information Review, vol. 30 no. 3
Type: Research Article
ISSN: 1468-4527

Keywords

Content available
Article
Publication date: 14 August 2007

232

Abstract

Details

Online Information Review, vol. 31 no. 4
Type: Research Article
ISSN: 1468-4527

Article
Publication date: 1 March 1974

Tom Schultheiss, Lorraine Hartline, Jean Mandeberg, Pam Petrich and Sue Stern

The following classified, annotated list of titles is intended to provide reference librarians with a current checklist of new reference books, and is designed to supplement the…

Abstract

The following classified, annotated list of titles is intended to provide reference librarians with a current checklist of new reference books, and is designed to supplement the RSR review column, “Recent Reference Books,” by Frances Neel Cheney. “Reference Books in Print” includes all additional books received prior to the inclusion deadline established for this issue. Appearance in this column does not preclude a later review in RSR. Publishers are urged to send a copy of all new reference books directly to RSR as soon as published, for immediate listing in “Reference Books in Print.” Reference books with imprints older than two years will not be included (with the exception of current reprints or older books newly acquired for distribution by another publisher). The column shall also occasionally include library science or other library related publications of other than a reference character.

Details

Reference Services Review, vol. 2 no. 3
Type: Research Article
ISSN: 0090-7324

1 – 9 of 9